Channel Selection in the Short-time Modulation Domain for Distant Speech Recognition; Comparison with the Envelope-variance Measure

نویسندگان

Ivan Himawan

Petr Motlicek

Sridha Sridharan

David Dean Dian Tjondronegoro

David Dean

Dian Tjondronegoro

چکیده

Automatic speech recognition from multiple distant microphones poses significant challenges because of noise and reverberations. The quality of speech acquisition may vary between microphones because of movements of speakers and channel distortions. This paper proposes a channel selection approach for selecting reliable channels based on selection criterion operating in the short-term modulation spectrum domain. The proposed approach quantifies the relative strength of speech from each microphone and speech obtained from beamforming modulations. The new technique is compared experimentally in the real reverb conditions in terms of perceptual evaluation of speech quality (PESQ) measures and word error rate (WER). Overall improvement in recognition rate is observed using delay-sum and superdirective beamformers compared to the case when the channel is selected randomly using circular microphone arrays.

برای دانلود رایگان متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

Channel selection in the short-time modulation domain for distant speech recognition

متن کامل

Robust Speech Recognition with Microphone Arrays in Multi-room Home Environments

This paper presents a set of exploratory experiments addressed to analyse and evaluate the performance of baseline speech processing components for distant voice command recognition applications in domestic environments. The analysis, conducted in a multi-channel multi-room scenario, showed the importance of adequate room detection and channel selection strategies to obtain acceptable performan...

متن کامل

A New Method for Speech Enhancement Based on Incoherent Model Learning in Wavelet Transform Domain

Quality of speech signal significantly reduces in the presence of environmental noise signals and leads to the imperfect performance of hearing aid devices, automatic speech recognition systems, and mobile phones. In this paper, the single channel speech enhancement of the corrupted signals by the additive noise signals is considered. A dictionary-based algorithm is proposed to train the speech...

متن کامل

On the potential of channel selection for recognition of reverberated speech with multiple microphones

The performance of ASR systems in a room environment with distant microphones is strongly affected by reverberation. As the degree of signal distortion varies among acoustic channels (i.e. microphones), the recognition accuracy can benefit from a proper channel selection. In this paper, we experimentally show that there exists a large margin for WER reduction by channel selection, and discuss s...

متن کامل

A Novel Frequency Domain Linearly Constrained Minimum Variance Filter for Speech Enhancement

A reliable speech enhancement method is important for speech applications as a pre-processing step to improve their overall performance. In this paper, we propose a novel frequency domain method for single channel speech enhancement. Conventional frequency domain methods usually neglect the correlation between neighboring time-frequency components of the signals. In the proposed method, we take...

متن کامل

ذخیره در منابع من

ذخیره در منابع من قبلا به منابع من ذحیره شده

{@ msg_add @}

با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

عنوان ژورنال:

دوره شماره

صفحات -

تاریخ انتشار 2015

Channel Selection in the Short-time Modulation Domain for Distant Speech Recognition; Comparison with the Envelope-variance Measure

نویسندگان

چکیده

منابع مشابه

Channel selection in the short-time modulation domain for distant speech recognition

Robust Speech Recognition with Microphone Arrays in Multi-room Home Environments

A New Method for Speech Enhancement Based on Incoherent Model Learning in Wavelet Transform Domain

On the potential of channel selection for recognition of reverberated speech with multiple microphones

A Novel Frequency Domain Linearly Constrained Minimum Variance Filter for Speech Enhancement

عنوان ژورنال:

اشتراک گذاری